Papers by Agustin Martin Picard
ConSim: Measuring Concept-Based Explanations’ Effectiveness with Automated Simulatability (2025.acl-long)
Copied to clipboard
| Challenge: | Existing evaluation metrics focus only on the quality of the induced space of possible concepts, neglecting the latter. |
| Approach: | They propose to use large language models as simulators to approximate the evaluation and report various analyses to make such approximations reliable. |
| Outcome: | The proposed framework allows for scalable and consistent evaluation across models and datasets. |